A method to locate protein coding sequences in DNA of prokaryotic systems
نویسندگان
چکیده
cDNA sequence data from E. coli phages, for which complete genome sequences are known, have been analysed, From this analysis thirteen triplets have been identified as markers to distinguish protein-coding frames from fortuitous open reading frames. The region of -18 to +18 nucleotides around ATG/GTG, has been analysed and used to identify initiator codons from internal ATG/GTG. With the aid of criteria defined above a method has been developed to locate protein coding sequences by a combination of 'gene search by signal' and 'gene search by content' approaches. Application of this method to prokaryotic systems including those which were not part of our data base indicates that it is quite accurate and general in nature.
منابع مشابه
تخمین مکان نواحی کدکننده پروتئین در توالی عددی DNA با استفاده پنجره با طول متغیر بر مبنای منحنی سه بعدی Z
In recent years, estimation of protein-coding regions in numerical deoxyribonucleic acid (DNA) sequences using signal processing tools has been a challenging issue in bioinformatics, owing to their 3-base periodicity. Several digital signal processing (DSP) tools have been applied in order to Identify the task and concentrated on assigning numerical values to the symbolic DNA sequence, then app...
متن کاملComputational Identification of Micro RNAs and Their Transcript Target(s) in Field Mustard (Brassica rapa L.)
Background: Micro RNAs (miRNAs) are a pivotal part of non-protein-coding endogenous small RNA molecules that regulate the genes involved in plant growth and development, and respond to biotic and abiotic environmental stresses posttranscriptionally.Objective: In the present study, we report the results of a systemic search for identifi cation of new miRNAs in B. rapa using homology-based ...
متن کاملInvestigation of Solvent Effect on CUA Codon Mutation: NMR Shielding Study
P53 is one of the gene that has important role in human cell cycle and in the human cancers too.Models of codon substitution make it possible to separate mutational biases in the DNA fromselective constraints on the protein, and offer a great advantage over amino acid models forunderstanding the evolutionary process of proteins and protein-coding DNA sequences. In thiswork, we investigated abou...
متن کاملبررسی مولکولی فعالیت پروتئولیتیکی پروتئین 3C در پاراکو ویروس انسانی تیپ یک
Background & Aim: Human parechovirus is a genus of picornaviruses which includes a number of important viruses of clinical significance. Recent sequence analysis suggests that human parechovirus type 1(HPEV-1) is distinct from other picornaviruses and lacks the motives believed to be involved in the protease function of 2A. The aim of this study was to analyse proteolytic activity of 3C pro...
متن کاملSegmentation of short human exons based on spectral features of double curves
This paper presents a new segmentation method based on spectral analysis to locate borders between short protein coding regions and non-coding regions. We formulate the innovative double curve representation of a DNA sequence and apply local three-codon measurement on the discrete Fourier spectral features at 1/3 frequency to identify short protein coding regions. The proposed spectral segmenta...
متن کاملذخیره در منابع من
با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید
برای دانلود متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید
ثبت ناماگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید
ورودعنوان ژورنال:
- Nucleic acids research
دوره 13 1 شماره
صفحات -
تاریخ انتشار 1985